deploying ml solution
Deploying ML solutions with low latency in Python
When we aim for better accuracies, sometimes we forget that the algorithms become more massive and slower. How do you deploy your solution? Which framework to use? Can you use Python for deploying my solution? If you are curious to solve these questions, join me in this talk to discover TensorRT and DeepStream and how they reduce your algorithm's latency and memory footprint. NVIDIA TensorRT is an SDK for high-performance deep learning inference.